Papers with speaker attribute prediction
Reinforced Multiple Instance Selection for Speaker Attribute Prediction (2024.naacl-long)
Copied to clipboard
Alireza Salkhordeh Ziabari, Ali Omrani, Parsa Hejabi, Preni Golazizian, Brendan Kennedy, Payam Piray, Morteza Dehghani
| Challenge: | Current methods for predicting speaker attributes take a speaker’s utterances as input and provide a prediction per speaker attribute. |
| Approach: | They propose a Multiple Instance Learning approach that uses Reinforcement Learning to predict speaker attributes using a set of utterances from social media posts and political ideologies from transcribed speeches. |
| Outcome: | The proposed approach outperforms existing methods on a range of related tasks including predicting speakers’ psychographics and demographics from social media posts and political ideologies from transcribed speeches. |
WenetSpeech-Wu: Datasets, Benchmarks, and Models for a Unified Chinese Wu Dialect Speech Processing Ecosystem (2026.findings-acl)
Copied to clipboard
Chengyou Wang, Mingchen Shao, Jingbin Hu, Zeyu Zhu, Hongfei Xue, Bingshen Mu, Xin Xu, Xingyi Duan, Binbin Zhang, Zhu Pengcheng, Chuang Ding, Xiaojun Zhang, Hui Bu, Lei Xie
| Challenge: | despite its linguistic significance, the Wu dialect of Chinese has long been hindered by the lack of large-scale speech data, standardized evaluation benchmarks, and publicly available models. |
| Approach: | They propose to use WenetSpeech-Wu as a large-scale, multi-dimensionally annotated open-source speech corpus for the Wu dialect of Chinese. |
| Outcome: | The proposed dataset includes 8,000 hours of speech data and strong open-source models . the proposed dataset is competitive and empirically validated . |